AIBase
Home
AI NEWS
AI Tools
GEO & AEO
MCP
AI Models
EN

AI News

View More

Vision-R1: Reinforcing Visual Localization with RL, Achieving 50% Performance Boost

Recently, a collaborative effort between the Institute of Automation, Chinese Academy of Sciences, and the CASIA-Zidong Taichu team introduced Vision-R1, a novel method leveraging R1-like reinforcement learning to significantly enhance visual localization capabilities. This approach achieved a 50% performance improvement in complex tasks such as object detection and visual localization, surpassing existing state-of-the-art (SOTA) models with over ten times the parameter scale. Currently, large vision-language models typically rely on "pre-training + supervised fine-tuning" to improve responsiveness to user instructions, but this...

10.6k 2 days ago
Vision-R1: Reinforcing Visual Localization with RL, Achieving 50% Performance Boost

Models

View More

DeepSeek-R1

Deepseek

DeepSeek-R1

$4

Input tokens/M

$16

Output tokens/M

32

Context Length

qwq-plus

Alibaba

qwq-plus

$1.6

Input tokens/M

$4

Output tokens/M

128

Context Length

DeepSeek-R1-Distill-Qwen-7B

Deepseek

DeepSeek-R1-Distill-Qwen-7B

$1

Input tokens/M

-

Output tokens/M

8

Context Length

ERNIE X1.1 Preview

Baidu

ERNIE X1.1 Preview

$1

Input tokens/M

$4

Output tokens/M

64

Context Length

o1

Openai

o1

$105

Input tokens/M

$420

Output tokens/M

200

Context Length

Qwen_v2.5_3b_Instruct

Alibaba

Qwen_v2.5_3b_Instruct

$1

Input tokens/M

-

Output tokens/M

32

Context Length

o1-mini

Openai

o1-mini

$21

Input tokens/M

$84

Output tokens/M

128

Context Length

o1-preview

Openai

o1-preview

$105

Input tokens/M

$420

Output tokens/M

128

Context Length

AIBase
Empowering the future, your artificial intelligence solution think tank
English简体中文繁體中文にほんご
FirendLinks:
AI Newsletters AI ToolsMCP ServersAI NewsAI MarketingLLM LeaderboardAI Ranking
© 2026AIBase
Business CooperationSite Map